DeepSWE v1.1 Leaderboard 2026: AI Models Ranked by Long-Horizon Engineering
Interactive DeepSWE v1.1 leaderboard with Claude Opus 5 at 74.0%, GPT-5.6 Sol at 72.7%, Grok 4.6 at 67.0%, and 24 models ranked by long-horizon software engineering ability. Updated August 14, 2026.
· 3.5K views
· Abdeladim Fadheli